AI Can Find the Bug. Who Is Going to Fix It?
Anthropic's new OSS Scanner can send maintainers unreviewed vulnerability reports from frontier models. The scanner is free. Triage, disclosure, patching, testing, and release work are not.
Anthropic's new OSS Scanner can send maintainers unreviewed vulnerability reports from frontier models. The scanner is free. Triage, disclosure, patching, testing, and release work are not.
OpenAI's Agents API turns the agent factory into rented infrastructure: hosted execution, durable sessions, subagents, browser use, and private MCP connections. The strategic product is no longer only the model. It is the control plane around the work.
NVIDIA's Open Agent Safety Platform moves agent controls outside the agent process, and its Sentry design moves the watchdog outside the host. That is the right architecture for a brake, but software availability is not proof that the whole system can contain a real incident.
OpenAI's move to wind down model access for SpaceX-owned Cursor is not just Musk-Altman drama. It is a reminder that coding agents depend on model suppliers, contracts, ownership, safety terms, and fallback plans.
Anthropic showed the primitives production agents need. Google is now packaging the same idea by industry with Gemini Enterprise for Legal and Financial Services. The useful signal is that agents are becoming governed workflow software: skills, connectors, permissions, citations, audit trails, and domain-specific control planes.
OpenAI's GPT-5.6 launch mattered. But ChatGPT Work was the clearer signal: agents were moving from answers to execution. OpenAI's new Enterprise Signals now gives the receipts: the firms pulling ahead are not just chatting more. They are giving agents context, tools, persistence, permissions, and review loops.
xAI's new Grok 4.5 is being sold as a coding and agentic-work model. The useful lesson is not the benchmark chart. It is the feedback loop between models, tools, and real work.